Papers by Soumya Smruti Mishra

4 papers
SLOT: Structuring the Output of Large Language Models (2025.emnlp-industry)

Copied to clipboard

Challenge: Structured outputs are essential for large language models (LLMs) but often deviate from predefined schemas hampering reliable application development.
Approach: They propose a model-agnostic approach that transforms unstructured LLM outputs into precise structured formats.
Outcome: The proposed model-agnostic approach transforms unstructured LLM outputs into precise structured formats.
A Systematic Survey of Automatic Prompt Optimization Techniques (2025.emnlp-main)

Copied to clipboard

Challenge: Recent advances in prompt engineering have created impediments for end users to adopt . however, prompt engineering remains an impedance due to rapid advances in models, tasks, and associated best practices.
Approach: They propose to define APO as a 5-part unifying framework and categorize all relevant works based on their salient features.
Outcome: The proposed framework aims to improve the performance of large language models on various tasks.
IPR: Intelligent Prompt Routing with User-Controlled Quality-Cost Trade-offs (2025.emnlp-industry)

Copied to clipboard

Challenge: Existing systems require users to manually select models or employ rigid routing rules that fail to capture the continuous spectrum of query complexity.
Approach: They propose a quality-constrained intelligent prompt routing framework that automatically selects optimal models based on predicted response quality and user-specified tolerance levels.
Outcome: The proposed framework achieves 43.9% cost reduction while maintaining quality parity with strongest model in the Claude family and processes requests with sub-150ms latency.
Reinforcement Learning for Self-Improving Agent with Skill Library (2026.acl-long)

Copied to clipboard

Challenge: Large Language Model (LLM)-based agents have demonstrated remarkable capabilities in complex reasoning and multi-turn interactions but struggle to continuously improve and adapt when deployed in new environments.
Approach: They propose a Reinforcement Learning-based approach to enhance agents’ self-improvement capabilities with a skill library.
Outcome: The proposed framework achieves 8.9% higher Scenario Goal Completion when applied to supervised-finetuned model with expert experience while requiring 26% fewer interaction steps and generating 59% fewer tokens.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations